Papers with Deep-Q Networks

1 papers
Feudal Reinforcement Learning for Dialogue Management in Large Domains (N18-2)

Copied to clipboard

Challenge: Reinforcement learning (RL) is a promising approach to model dialogue policy optimisation but fails to scale to large domains due to the curse of dimensionality.
Approach: They propose a novel approach to dialogue policy optimisation using reinforcement learning . they propose to decompose the decision into two steps using a domain ontology .
Outcome: The proposed architecture outperforms state-of-the-art in several dialogue domains without any additional reward signal.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations